Repository navigation
support transformers 5.14.1 - #2024
Merged
Merged
Conversation
jayhenry
approved these changes
Aug 18, 2026
longboat2010
added a commit
to Ascend-SHL-SACT/xtuner
that referenced
this pull request
Aug 20, 2026
…Split) (#30) * [ci] change lmdeploy version and related packages (InternLM#1996) * [ci] update vllm case in gpu * fir error on rl vllm case * fix script error * fix error * [Fix] Fix Ray generate concurrency group metadata (InternLM#2005) * Fix Ray generate concurrency group metadata * fix docs build error * [Fix] Avoid scanning Transformers lazy module during test collection (InternLM#2008) * test: avoid scanning Transformers lazy module * test: call dense decoder with keyword arguments * test: preserve Qwen3.5 vision interpolation dtype * Fix custom preprocess and postprocess in judger pools (InternLM#2000) * [ci] Add Intern-S2-Preview RL ETE coverage and user examples (InternLM#2010) * [CI] Expand Claude review guidance (InternLM#2012) * [CI] Use latest Claude Opus model (InternLM#2013) * [Fix] discard expired states when tail batch is disabled (InternLM#2006) * discard expired states when tail batch is disabled * delete bind function and add retryable attr * [GLM-5.2] Preserve HF config compatibility for vLLM (InternLM#1998) * [GLM-5.2] Preserve HF export compatibility fields Keep legacy routing metadata required by vLLM and trim unused RoPE defaults from exported configs. * [Testing] Add HF config export contract checker * [Skills] Handle missing HF config exporters * [Fix] Supervise GLM-5.2 assistant stop tokens (InternLM#1997) * [Fix] Supervise GLM-5.2 assistant stop tokens * [SKILL] Add chat-template audit and implementation skill Co-authored-by: Cursor <cursoragent@cursor.com> * [SKILL] Remove Python environment instruction --------- Co-authored-by: Cursor <cursoragent@cursor.com> * [CI] Improve Claude action completion and tool access (InternLM#2015) * [CI] Improve Claude review completion * [CI] Expose Claude review turn budget * [CI] Expand sandboxed Claude review tools * [CI] Rely on isolated review runner * [CI] Expand Claude comment capabilities * [CI] Define Claude review section order * [CI] Add Claude test review guidance * [Fix] Add GLM-5.2 MuonSplit and AdamW-only gradient clipping (InternLM#2001) * Fix GLM-5.2 Muon splitting and gradient clipping * Fix MuonSplit parameter typing * Refactor gradient clipping policy * [CI] Split Claude review analysis and publication (InternLM#2016) * [CI] Split Claude review into analysis and publish phases * fix(ci): publish Claude review summary deterministically * fix(ci): scope Claude review to prepared PR diff * fix(ci): harden Claude review credentials * fix(ci): prepare deterministic Claude review bundle * refactor(ci): simplify Claude review workflow * [CI] Streamline Claude review output and publication (InternLM#2019) * [CI] Fix Claude review summary publication Allow the StructuredOutput tool required by --json-schema instead of denying every tool, and give the summary phase a second turn. Limit reported findings to Warning severity and above to keep reviews concise. * [CI] Streamline Claude review prompt and deduplication Limit review output to supported Warning+ findings, preserve the summary contract, and add a normalized discussion index for efficient deduplication. * [CI] Refine Claude review comment format Classify inline findings, keep routine comments within 150 characters, and state the required summary section order. * add qwen3.5 test cases about ep_size (InternLM#2011) * add ep case * update config * add rl mtp config * debug * update step * debug * debug * add ci debug env * update config * update description * [ci] optimize ete false positive (InternLM#2028) * Reduce ETE false positives from tight KL/time thresholds and offline resume first checks. Loosen qwen3-5 VL RL mismatch_k3_kl and qwen3-rl-lmdeploy KL/time budgets, and compare only the first-run tracker prefix when phase=first sees a merged offline file. * Check mismatch_k3_kl per-step value < 0.001 instead of baseline drift. Add method=value for RL metric bounds and apply it to all K3 KL checks so ETE no longer fails on run-to-run abs diffs. * Align qwen3-5-rl-vl-lmdeploy-mtp-ep K3 KL check to per-step value < 0.001. * update * Harden ETE checks: fix RL resume meta, drop SFT round(,2), slowdown-only time. Resolve colocate RL meta for update_meta, compare SFT relative error without two-decimal rounding, and only penalize time/step regressions for s2-preview. * drop faild (InternLM#1945) Improve agent rollout failure handling * [Feat] Add off-policy masking for partial rollouts (InternLM#2003) * Add token staleness masking during batch take * Add lifecycle-aware sampling for token-expired groups * Refine expired rollout sampling and tail-batch semantics * Refine staleness lifecycle for expired groups * fix claude comments * [Fix] fix RL MTP config handling for non-compose models (InternLM#2031) Fix RL MTP config handling for non-compose models * support transformers 5.14.1 (InternLM#2024) * support transformers 5.14.1 * update * [ci] Optimize ete case, relax sft metric after dropping round (InternLM#2034) * [ci] relax SFT metric thresholds after dropping round(,2) Align loss and text_tokens tolerances with observed run-to-run drift so ETE baseline checks match unrounded relative error semantics. Co-authored-by: Cursor <cursoragent@cursor.com> * [ci] use absolute diff for text_tokens in SFT metric checks Add per-metric comparison method support and compare token counts with exact-match threshold 0 while keeping relative drift checks for loss metrics. Co-authored-by: Cursor <cursoragent@cursor.com> --------- Co-authored-by: JuliaLin <julialin@JuliaLindeMacBook-Pro.local> Co-authored-by: Cursor <cursoragent@cursor.com> * [Fix] Preserve concurrent trace sessions in disaggregated RL (InternLM#2021) * [Fix] Preserve concurrent rollout trace sessions * [Fix] Address trace-session lifecycle review * refactor(rl): centralize terminal trace cleanup * refactor(rl): use non-retryable cleanup terminology * style(rl): apply docformatter * fix(rl): invalidate trace store cache without Ray --------- Co-authored-by: zhulinJulia24 <145004780+zhulinJulia24@users.noreply.github.com> Co-authored-by: kkscilife <126147887+kkscilife@users.noreply.github.com> Co-authored-by: Yanhui Duan <dyh10280@163.com> Co-authored-by: Penghao Zhao <henryzhao1989@163.com> Co-authored-by: Cursor <cursoragent@cursor.com> Co-authored-by: 纪焘 (Tao Ji) <taoji.cs@gmail.com> Co-authored-by: liukuikun <24622904+Harold-lkk@users.noreply.github.com> Co-authored-by: PengchengShi00 <146822991+PengchengShi00@users.noreply.github.com> Co-authored-by: Haian Huang(深度眸) <1286304229@qq.com> Co-authored-by: JuliaLin <julialin@JuliaLindeMacBook-Pro.local> Co-authored-by: matrix72 <60974665+matrix72c@users.noreply.github.com>
This file contains hidden or bidirectional Unicode text that may be interpreted or compiled differently than what appears below. To review, open the file in an editor that reveals hidden Unicode characters.
Learn more about bidirectional Unicode characters
Sign up for free
to join this conversation on GitHub.
Already have an account?
Sign in to comment
Add this suggestion to a batch that can be applied as a single commit.This suggestion is invalid because no changes were made to the code.Suggestions cannot be applied while the pull request is closed.Suggestions cannot be applied while viewing a subset of changes.Only one suggestion per line can be applied in a batch.Add this suggestion to a batch that can be applied as a single commit.Applying suggestions on deleted lines is not supported.You must change the existing code in this line in order to create a valid suggestion.Outdated suggestions cannot be applied.This suggestion has been applied or marked resolved.Suggestions cannot be applied from pending reviews.Suggestions cannot be applied on multi-line comments.Suggestions cannot be applied while the pull request is queued to merge.Suggestion cannot be applied right now. Please check back later.
Uh oh!
There was an error while loading. Please reload this page.